Investigating Open Source Project Success: A Data Mining Approach to Model Formulation, Validation and Testing
نویسندگان
چکیده
This paper demonstrates the use of Data Mining (DM) techniques in exploratory research. A robust model for identifying the factors that explain the success of Open Source Software (OSS) projects is created, validated and tested. The predictive modeling techniques of Logistic Regression (LR), Decision Trees (DT) and Neural Networks (NN) are used together in this analysis. Using Text Mining results in the predictive modeling process strengthens the model. SAS ® Enterprise Miner and SAS ® Text Miner are used in this research.
منابع مشابه
Antecedents of open source software defects: A data mining approach to model formulation, validation and testing
This paper develops tests and validates a model for the antecedents of open source software (OSS) defects, using Data and Text Mining. The public archives of OSS projects are used to access historical data on over 5,000 active and mature OSS projects. Using domain knowledge and exploratory analysis, a wide range of variables is identified from the process, product, resource, and end-user charac...
متن کاملPredicting OSS Development Success: A Data Mining Approach
Open Source Software (OSS) has reached new levels of sophistication and acceptance by users and commercial software vendors. This research creates tests and validates a model for predicting successful development of OSS projects. Widely available archival data was used for OSS projects from Sourceforge. net. The data is analyzed with multiple Data Mining techniques. Initially three competing mo...
متن کاملRobust production scheduling in open-pit mining under uncertainty: a box counterpart approach
Open-Pit Production Scheduling (OPPS) problem focuses on determining a block sequencing and scheduling to maximize Net Present Value (NPV) of the venture under constraints. The scheduling model is critically sensitive to the economic value volatility of block, block weight, and operational capacity. In order to deal with the OPPS uncertainties, various approaches can be recommended. Robust opti...
متن کاملSpatial modelling of zonality elements based on compositional nature of geochemical data using geostatistical approach: a case study of Baghqloom area, Iran
Due to the existence of a constant sum of constraints, the geochemical data is presented as the compositional data that has a closed number system. A closed number system is a dataset that includes several variables. The summation value of variables is constant, being equal to one. By calculating the correlation coefficient of a closed number system and comparing it with an open number system, ...
متن کاملA New Cost Model for Estimation of Open Pit Copper Mine Capital Expenditure
One of the most important issues in all stages of mining study is capital cost estimation. Determination of capital expenditure is a challenging issue for mine designers. In recent decade, quite a few number of studies have focused on proposing estimation models to predict mining capital cost. However, these efforts have not achieved to a predictor model with reliable range of error. Both of ov...
متن کامل